Papers with speech community

7 papers
Agent-based Modeling of Language Change in a Small-world Network (2024.lrec-main)

Copied to clipboard

Challenge: linguists have been studying how external aspects of the linguistic system impact the grammatical system of a language.
Approach: They propose to use agent-based modeling and simulation to study language change in a speech community using Zachary’s karate club network.
Outcome: The proposed model shows that the centrality of each agent in the network, interpreted as social prestige, appears to be a factor influencing change.
Dim Wihl Gat Tun: The Case for Linguistic Expertise in NLP for Under-Documented Languages (2022.findings-acl)

Copied to clipboard

Challenge: Recent progress in NLP is driven by pretrained models leveraging massive datasets.
Approach: They argue that IGT data can be leveraged provided target language expertise is available and that it can be used to create effective models.
Outcome: The proposed model can be leveraged provided that target language expertise is available.
Generative Spoken Language Model based on continuous word-sized audio tokens (2023.emnlp-main)

Copied to clipboard

Challenge: Text-based language models outperform character-based models, but speech inputs are 20ms or 40ms-long discrete units.
Approach: They propose a generative language model based on word-size continuous audio tokens . they replace lookup table for lexical types with a Lexical Embedding function .
Outcome: The proposed model is five times more memory efficient than discrete unit GSLMs and is phonetically and semantically interpretable.
Far-Field Speaker Recognition Benchmark Derived From The DiPCo Corpus (2022.lrec-1)

Copied to clipboard

Challenge: Using a publicly-available corpus, we propose a far-field speaker verification benchmark.
Approach: They propose a far-field speaker verification benchmark derived from the publicly available DiPCo corpus.
Outcome: The proposed tasks are very challenging and hope to inspire the speech community to develop new methods and systems for this challenging domain.
Speech Resources in the Tamasheq Language (2022.lrec-1)

Copied to clipboard

Challenge: In this paper, we present two datasets for Tamasheq, a developing language mainly spoken in Mali and Niger . we share unlabeled audio data in five languages: french, Fulfulde, Hausa, Tamaheq and Zarma .
Approach: They present two datasets for Tamasheq, a developing language mainly spoken in Mali and Niger.
Outcome: The proposed datasets are used in the IWSLT 2022 low-resource speech translation track . they consist of radio recordings from daily broadcast news in Niger and Mali .
Building Open Javanese and Sundanese Corpora for Multilingual Text-to-Speech (L18-1)

Copied to clipboard

Challenge: Using multi-speaker text-to-speech systems, we build systems for Javanese and Sundanese . progress in this direction is difficult because languages in the long tail of the distribution of the majority of the world's languages lack adequate linguistic resources .
Approach: They present multi-speaker text-to-speech corpora for Javanese and Sundanese . they use mixed-gender recordings to build multi-language text-based systems .
Outcome: The proposed multi-speaker text-to-speech systems outperform the systems constructed from a single language.
Making “fetch” happen: The influence of social and linguistic context on nonstandard word growth and decline (D18-1)

Copied to clipboard

Challenge: In an online community, new words come and go, but language change is shaped and constrained by the grammatical system in which it takes part.
Approach: They analysed the frequency of non-standard words in reddit to determine their impact on language change.
Outcome: The results show that language change is shaped and constrained by the grammatical system in which it takes place.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations